Voice Activity Detection

نویسندگان

  • Jonathan Kola
  • Carol Espy-Wilson
  • Tarun Pruthi
چکیده

• Voice activity detection is the process by which algorithms called Voice Activity Detectors (VADs) are able to distinguish regions that contain speech from regions that do not contain speech in an audio signal • Several features distinguish speech from nonspeech, however, where the speech signal is corrupted by background noise it becomes more and more difficult to characterize these features and make a decision

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

A New Algorithm for Voice Activity Detection Based on Wavelet Packets (RESEARCH NOTE)

Speech constitutes much of the communicated information; most other perceived audio signals do not carry nearly as much information. Indeed, much of the non-speech signals maybe classified as ‘noise’ in human communication. The process of separating conversational speech and noise is termed voice activity detection (VAD). This paper describes a new approach to VAD which is based on the Wavelet ...

متن کامل

An End-to-End Architecture for Keyword Spotting and Voice Activity Detection

We propose a single neural network architecture for two tasks: on-line keyword spotting and voice activity detection. We develop novel inference algorithms for an end-to-end Recurrent Neural Network trained with the Connectionist Temporal Classification loss function which allow our model to achieve high accuracy on both keyword spotting and voice activity detection without retraining. In contr...

متن کامل

Voice Activity Detection based on Inverse Normalized Noise Likelihood Estimation

In this paper we develop a voice activity detection algorithm based on the likelihood that only noise is present in the current signal frame. For this we exploit the fact that the Fourier coefficients of most noise processes can be modeled as statistically independent Gaussian random variables. We also give an overview of different voice activity detectors previously described in the literature...

متن کامل

Voice Activity Detection Using Generalized Gamma Distribution

In this work, we model speech samples with a two-sided generalized Gamma distribution and evaluate its efficiency for voice activity detection. Using a computationally inexpensive maximum likelihood approach, we employ the Bayesian Information Criterion for identifying the phoneme boundaries in noisy speech.

متن کامل

A Wavelet-Based Voice Activity Detection Algorithm in Variable-Level Noise Environment

In this paper, a novel entropy-based voice activity detection (VAD) algorithm is presented in variable-level noise environment. Since the frequency energy of different types of noise focuses on different frequency subband, the effect of corrupted noise on each frequency subband is different. It is found that the seriously obscured frequency subbands have little word signal information left, and...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 2011